Papers with text-image pairs
Multimodal Aspect-Based Sentiment Analysis under Conditional Relation (2025.coling-main)
Copied to clipboard
| Challenge: | Existing methods to analyze social media sentiments rely on image-based aspects. |
| Approach: | They propose a multi-task framework to extract aspect terms from text-image pairs and identify their sentiments. |
| Outcome: | The proposed framework outperforms existing methods on a text-image dataset. |
AoM: Detecting Aspect-oriented Information for Multimodal Aspect-Based Sentiment Analysis (2023.findings-acl)
Copied to clipboard
| Challenge: | Existing methods to extract aspects from text-image pairs and recognize their sentiments are noisy and coarsely establishing image-aspect alignment will interfere with aspect-relevant semantic and sentiment information. |
| Approach: | They propose an Aspect-oriented method to detect aspect-relevant semantic and sentiment information by selecting textual tokens and image blocks that are semantically related to the aspects. |
| Outcome: | The proposed method is superior to existing methods in the field of sentiment analysis. |
Incongruity-aware Tension Field Network for Multi-modal Sarcasm Detection (2025.acl-long)
Copied to clipboard
| Challenge: | Multi-modal sarcasm detection (MSD) identifies sarcasm and accurately understands users’ real attitudes from text-image pairs. |
| Approach: | They propose to use incongruity-aware tension field network to extract effective text-image feature pairs in fact and sentiment perspectives and construct a fact/sentiment tension field with discrepancy metrics to capture contextual tone and polarized inconcongruities. |
| Outcome: | The proposed method achieves state-of-the-art performance surpassing LLaVA1.5-7B with only 17.3M trainable parameters, demonstrating its optimal performance-efficiency in multi-modal sarcasm detection tasks. |